首页> 外文OA文献 >Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

【2h】

Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

机译：知道何时看：通过Visual sentinel for Image的自适应注意字幕

代理获取

本网站仅为用户提供外文OA文献查询和代理获取服务，本网站没有原文。下单后我们将采用程序或人工为您竭诚获取高质量的原文，但由于OA文献来源多样且变更频繁，仍可能出现获取不到、文献不完整或与标题不符等情况，如果获取不到我们将提供退款服务。请知悉。

页面导航

摘要
著录项
相似文献
相关主题

摘要

Attention-based neural encoder-decoder frameworks have been widely adoptedfor image captioning. Most methods force visual attention to be active forevery generated word. However, the decoder likely requires little to no visualinformation from the image to predict non-visual words such as "the" and "of".Other words that may seem visual can often be predicted reliably just from thelanguage model e.g., "sign" after "behind a red stop" or "phone" following"talking on a cell". In this paper, we propose a novel adaptive attention modelwith a visual sentinel. At each time step, our model decides whether to attendto the image (and if so, to which regions) or to the visual sentinel. The modeldecides whether to attend to the image and where, in order to extractmeaningful information for sequential word generation. We test our method onthe COCO image captioning 2015 challenge dataset and Flickr30K. Our approachsets the new state-of-the-art by a significant margin.

机译：基于注意力的神经编码器-解码器框架已广泛用于图像字幕。大多数方法会迫使视觉注意力永远活跃在每个生成的单词上。但是，解码器可能几乎不需要图像的视觉信息即可预测诸如“ the”和“ of”之类的非视觉单词。看起来看似视觉的其他单词通常可以仅根据语言模型可靠地进行预测，例如，“ sign”之后“在单元格上交谈”之后的“红色停止点后面”或“电话”。在本文中，我们提出了一种具有视觉前哨的新型自适应注意力模型。在每个时间步长，我们的模型都决定是关注图像（如果是，则关注哪个区域）还是视觉标记。该模型决定是否关注图像以及在哪里关注，以便提取有意义的信息以进行顺序的单词生成。我们在COCO图片字幕2015挑战数据集和Flickr30K上测试了我们的方法。我们的方法极大地设置了新的最新技术。

著录项

作者
Lu, Jiasen; Xiong, Caiming; Parikh, Devi; Socher, Richard;
展开▼
作者单位

展开▼
年度 2017
总页数
原文格式 PDF
正文语种
中图分类

相似文献

外文文献
中文文献
专利

1. Monthly composites from Sentinel-1 and Sentinel-2 images for regional major crop mapping with Google Earth Engine [J] . LUO Chong, LIU Huan-jun, LU Lii-ping, 农业科学学报：英文版 . 2021,第007期
2. Monthly composites from Sentinel-1 and Sentinel-2 images for regional major crop mapping with Google Earth Engine [J] . LUO Chong, LIU Huan-jun, LU Lü-ping, 农业科学学报（英文版） . 2021,第007期
3. Decoding brain responses to pixelized images in the primary visual cortex: implications for visual cortical prostheses [J] . Bing-bing Guo, Xiao-lin Zheng, Zhen-gang Lu, 中国神经再生研究（英文版） . 2015,第010期
4. Winter wheat yield estimation based on assimilated Sentinel-2 images with the CERES-Wheat model [J] . LIU Zheng-chun, WANG Chao, Bl Ru-tian, 农业科学学报：英文版 . 2021,第007期
5. Winter wheat yield estimation based on assimilated Sentinel-2 images with the CERES-Wheat model [J] . LIU Zheng-chun, WANG Chao, BI Ru-tian, 农业科学学报（英文版） . 2021,第007期
6. Hierarchical LSTMs with Adaptive Attention for Visual Captioning [J] . Gao Lianli, Li Xiangpeng, Song Jingkuan, IEEE Transactions on Pattern Analysis and Machine Intelligence . 2020,第5期

机译：具有自适应关注的分层LSTMS对视觉标题
7. Clothes image caption generation with attribute detection and visual attention model [J] . Li Xianrui, Ye Zhiling, Zhang Zhao, Pattern recognition letters . 2021,第Jana期

机译：衣服图像标题生成，具有属性检测和视觉注意模型
8. Exploring region relationships implicitly: Image captioning with visual relationship attention [J] . Zhang Zongjian, Wu Qiang, Wang Yang, Image and Vision Computing . 2021,第May期

机译：隐含地探索区域关系：具有视觉关系的图像标题
9. Knowing When to Look: Adaptive Attention via a Visual Sentinel for Image Captioning [C] . Jiasen Lu, Caiming Xiong, Devi Parikh, IEEE Conference on Computer Vision and Pattern Recognition . 2017

机译：知道何时看：通过视觉前哨进行自适应注意力以进行图像字幕
10. Arabic Image Captioning Using Deep Learning with Attention [D] . Sabri, Sabri Monaf. 2021

机译：使用深入学习的阿拉伯语图像标题
11. Social Image Captioning: Exploring Visual Attention and User Attention [O] . Leiquan Wang, Xiaoliang Chu, Weishan Zhang, 2018

机译：社交图像字幕：探索视觉注意力和用户注意力
12. Social Image Captioning: Exploring Visual Attention and User Attention [O] . Leiquan Wang, Xiaoliang Chu, Weishan Zhang, 2018

机译：社会形象标题：探索视觉关注和用户注意力

Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

摘要

著录项

相似文献

相关主题

期刊订阅